Papers by Agustin Martin Picard

1 papers
ConSim: Measuring Concept-Based Explanations’ Effectiveness with Automated Simulatability (2025.acl-long)

Copied to clipboard

Challenge: Existing evaluation metrics focus only on the quality of the induced space of possible concepts, neglecting the latter.
Approach: They propose to use large language models as simulators to approximate the evaluation and report various analyses to make such approximations reliable.
Outcome: The proposed framework allows for scalable and consistent evaluation across models and datasets.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations